Foundry Local

Foundry Local enables local execution of Small Language Models using the hardware on your device.

Details
RagIQ RuntimeEngine

Lightweight, highly optimized CPU runtime for GGUF models and embeddings.

Details
dmr

Standalone Docker Model Runner - run local AI models with no Docker Desktop or Engine required.

Details
BitLlama Desktop

Desktop GUI for BitLlama LLM inference engine with Soul learning and model management

Details